LTX2: Efficient Joint Audio-Visual Foundation Model
LTX2 is a foundation model for text-to-audio-video (T2AV) and text-image-to-audio-video (TI2AV) with the input image being the first frame of video.
LTX2 is a foundation model for text-to-audio-video (T2AV) and text-image-to-audio-video (TI2AV) with the input image being the first frame of video.
Date: 2026-09-27 Sun